Papers with MGSM benchmark

1 papers
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning (2026.findings-eacl)

Copied to clipboard

Challenge: Recent studies have focused on improving reasoning ability in English models, with multilingual models receiving comparatively little attention.
Approach: They propose a framework that ranks candidate reasoning traces across languages rather than within a single language.
Outcome: The proposed framework improves accuracy by up to 10 points in English compared to using reward modeling within a single language.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations